Search Results for "reinforcement learning"

NVIDIA Unveils Llama 3.1-Nemotron-70B-Reward to Enhance AI Alignment with Human Preferences

NVIDIA Unveils Llama 3.1-Nemotron-70B-Reward to Enhance AI Alignment with Human Preferences

NVIDIA introduces Llama 3.1-Nemotron-70B-Reward, a leading reward model that improves AI alignment with human preferences using RLHF, topping the RewardBench leaderboard.

Exploring Open Source Reinforcement Learning Libraries for LLMs

Exploring Open Source Reinforcement Learning Libraries for LLMs

An in-depth analysis of leading open-source reinforcement learning libraries for large language models, comparing frameworks like TRL, Verl, and RAGEN.

DeepSWE: Revolutionizing Coding Agents with Open-Source Reinforcement Learning

DeepSWE: Revolutionizing Coding Agents with Open-Source Reinforcement Learning

DeepSWE-Preview, an advanced coding agent, sets new benchmarks in open-source AI with a 59% success rate on SWE-Bench-Verified, showcasing state-of-the-art performance using reinforcement learning.

NVIDIA NeMo-RL Utilizes GRPO for Advanced Reinforcement Learning

NVIDIA NeMo-RL Utilizes GRPO for Advanced Reinforcement Learning

NVIDIA introduces NeMo-RL, an open-source library for reinforcement learning, enabling scalable training with GRPO and integration with Hugging Face models.

NVIDIA's ProRL v2 Advances LLM Reinforcement Learning with Extended Training

NVIDIA's ProRL v2 Advances LLM Reinforcement Learning with Extended Training

NVIDIA unveils ProRL v2, a significant leap in reinforcement learning for large language models (LLMs), enhancing performance through extended training and innovative algorithms.

TorchForge RL Pipelines Now Operable on Together AI's Cloud

TorchForge RL Pipelines Now Operable on Together AI's Cloud

Together AI introduces TorchForge RL pipelines on its cloud platform, enhancing distributed training and sandboxed environments with a BlackJack training demo.

Leveraging Reinforcement Learning for Scientific AI Agents

Leveraging Reinforcement Learning for Scientific AI Agents

Explore how reinforcement learning enhances scientific AI agents, reducing the burden of repetitive tasks and fostering innovation, as detailed by NVIDIA.

NVIDIA Unveils AI Agent Training Method Using Synthetic Data and GRPO

NVIDIA Unveils AI Agent Training Method Using Synthetic Data and GRPO

NVIDIA's new approach combines synthetic data generation with reinforcement learning to train CLI agents on a single GPU, cutting training time from months to days.

SkyRL Adds Vision-Language RL Support for Multimodal Models

SkyRL Adds Vision-Language RL Support for Multimodal Models

SkyRL introduces vision-language reinforcement learning, enabling scalable training for multimodal tasks. Learn how this impacts AI development.

NVIDIA Explores RLVR for AI Agents with Nemotron 3 Super

NVIDIA Explores RLVR for AI Agents with Nemotron 3 Super

NVIDIA showcases Nemotron 3 Super and RLVR techniques to improve AI agents' domain-specific workflows, pushing reinforcement learning's practical limits.

NVIDIA Vera Rubin Optimizes AI Post-Training Efficiency

NVIDIA Vera Rubin Optimizes AI Post-Training Efficiency

NVIDIA’s Vera Rubin platform sets a new benchmark for AI post-training, maximizing intelligence per dollar in the agentic AI era.

Machine Learning Network Fetch.ai Shares Vision of Interoperable Blockchains

Machine Learning Network Fetch.ai Shares Vision of Interoperable Blockchains

Fetch.ai has a vision: to bring machine learning services to every ledger and blockchain.

Trending topics